AIBase
Home
AI NEWS
AI Tools
GEO & AEO
MCP
AI Models
EN

AI News

View More

Xiaomi Open Sources Its First Native End-to-End Speech Large Model Xiaomi-MiMo-Audio

On September 19, Xiaomi announced the open source of its first native end-to-end speech large model, Xiaomi-MiMo-Audio, which marks a major breakthrough in the field of speech technology. Five years ago, the emergence of GPT-3 opened a new era for general artificial intelligence (AGI) in language, but the speech field has long been constrained by reliance on large-scale annotated data, making it difficult to achieve similar few-shot generalization capabilities as language models. Now, Xiaomi's Xiaomi-MiMo-Audio model is based on innovative pre-training

13.4k 3 days ago
Xiaomi Open Sources Its First Native End-to-End Speech Large Model Xiaomi-MiMo-Audio

Models

View More

Gemini 2.0 Flash

Google

Gemini 2.0 Flash

$0.7

Input tokens/M

$2.8

Output tokens/M

1k

Context Length

Gemini 2.5 Flash

Google

Gemini 2.5 Flash

$2.1

Input tokens/M

$17.5

Output tokens/M

1k

Context Length

GPT-5 nano

Openai

GPT-5 nano

$0.35

Input tokens/M

$2.8

Output tokens/M

400

Context Length

GPT OSS 120B

Openai

GPT OSS 120B

$0.63

Input tokens/M

$3.15

Output tokens/M

131

Context Length

Gemma 3n E4B

Google

Gemma 3n E4B

-

Input tokens/M

-

Output tokens/M

-

Context Length

Gemma 3n E2B Instructed

Google

Gemma 3n E2B Instructed

-

Input tokens/M

-

Output tokens/M

-

Context Length

Gemma 3n E4B Instructed LiteRT Preview

Google

Gemma 3n E4B Instructed LiteRT Preview

-

Input tokens/M

-

Output tokens/M

-

Context Length

Gemma 3n E4B Instructed

Google

Gemma 3n E4B Instructed

$140

Input tokens/M

$280

Output tokens/M

32

Context Length

AIBase
Empowering the future, your artificial intelligence solution think tank
English简体中文繁體中文にほんご
FirendLinks:
AI Newsletters AI ToolsMCP ServersAI NewsAI MarketingLLM LeaderboardAI Ranking
© 2026AIBase
Business CooperationSite Map